About this role.
This is a Principal DevOps Engineer role at Zartis, a global AI transformation consulting partner, working on a founding-stage initiative to build an internal agentic platform for a large digital marketplace group. The role involves designing secure agent runtimes, building evaluation pipelines, implementing observability and cost control, and creating reusable scaffolding for multi-cloud, multi-stack environments. The ideal candidate has deep experience in developer platforms, secure execution environments, and MCP-style integrations. This is a high-impact, high-autonomy position requiring strong SRE and instrumentation skills, with a focus on enabling AI-driven workflows across multiple marketplace products.
Role DNA
A quick view of the complexity, pace, ownership and collaboration implied by the job description.
Job Complexity
5/5Pace & Pressure
5/5Autonomy Level
5/5Communication Load
3/5Salary analysis
Estimated compensation compared with the broader US market for similar roles.
Core skills
Skills and capabilities most closely associated with this opportunity.
Cover letter sample
Dear Hiring Team,
I am writing to express my strong interest in the Principal DevOps Engineer role at Zartis. With over a decade of experience building and scaling internal developer platforms, I am excited by the opportunity to architect the agentic platform from its inception. My background includes designing secure execution environments with sandboxing, RBAC, and audit logging, as well as integrating MCP-style tool gateways across diverse systems.
At my previous company, I led the creation of an evaluation pipeline with golden tasks and release gates that reduced agent workflow failures by 40%. I have an instrumentation-first mindset, building observability and cost-control systems that give stakeholders real-time visibility into performance and expenses. I thrive in founding-stage environments where I can define the architecture and drive it to production.
I am particularly drawn to Zartis's focus on AI-driven platforms and the opportunity to work across a multi-cloud, multi-stack estate. I am confident that my technical expertise and collaborative approach will help accelerate the adoption of agentic engineering across the marketplace group. I look forward to discussing how I can contribute to your team.
Sincerely,
[Your Name]
Sample interview questions
I led the creation of an internal platform that provided self-service CI/CD pipelines, environment management, and observability. Key decisions included using Kubernetes as the substrate, Terraform for infrastructure, and a service catalog with RBAC. To drive adoption, we partnered with early adopter teams, provided extensive documentation and workshops, and iterated based on feedback. We also set up SLAs and on-call rotations to build trust.
I would use sandboxing via gVisor or Firecracker micro-VMs for isolation. Implement network policies to restrict egress, secrets management with Vault, and RBAC for fine-grained access control. Audit logging all actions. For approval flows, require human review for sensitive operations. Also, rate limiting and resource quotas to prevent abuse.
I would define golden tasks with expected outputs, use automated graders (e.g., LLM-as-judge) and regression tests. Metrics include success rate, latency, cost per task, and failure modes. To keep costs low, use spot instances for evaluation, cache results, and run only on code changes. Also, provide a dashboard for teams to see pass/fail rates.
I use OpenTelemetry for traces and metrics, Prometheus for monitoring, and Grafana for dashboards. Centralized logging with Loki. For cost attribution, I label resources and use cloud cost APIs. I implement structured logging and distributed tracing across services. For agent workflows, I track token usage and step-level latency.
At a previous job, we needed to deliver a new self-service feature quickly. I proposed launching with a limited feature set and a manual approval gate for safety. We communicated openly with users about the trade-offs. Over time, we automated approvals and added more capabilities. This allowed us to iterate fast without compromising production stability.
Zartis is a global AI transformation and technology consulting partner where talented engineers and technologists work on cutting-edge innovation. We partner with ambitious organizations to design, build, and scale technology solutions that deliver real impact.
Our teams bring deep expertise in AI-driven platforms, secure API architectures, and cloud-native engineering. You will work on meaningful projects that accelerate the adoption of advanced technologies, from strategy and discovery through to full product delivery, helping turn complex challenges into measurable outcomes.
With engineering hubs across EMEA and LATAM, and long-term partnerships in financial services, healthcare and life sciences, and energy and climate, we offer opportunities to work on projects that truly matter. Here, you will not just build technology — you will drive business impact and grow your career alongside industry leaders.
We are looking for a Principal DevOps Engineer to work on a project in the marketplace sector.
The project:
We are looking for a talented Principal DevOps Engineer to join a founding-stage initiative at one of Europe’s largest digital-native groups, building the internal agentic platform that enables safe, observable, and reusable AI-driven workflows at marketplace scale.
You will be part of a founding platform cell at the heart of an AI House, the central engine for making agentic engineering the default across all marketplace teams. The platform does not exist yet: you help define what it is. You will work across secure agent runtimes, tool gateways, evaluation pipelines, observability systems, and reusable scaffolding, all built for a heterogeneous, multi-cloud, multi-stack estate spanning products like Leboncoin, Kleinanzeigen, Marktplaats, Subito, Mobile.de, and more.
Our teammates come from a variety of backgrounds and we are committed to building an inclusive culture based on trust and innovation.
What you will do:
Build and operate secure agent runtimes with sandboxing, runtime isolation, network policies, secrets management, RBAC, and approval flows that make agents a credible part of the infrastructure.
Design and maintain the integration surface with MCP-style adapters and gateways that let agents act on source control, CI/CD, ticketing, documentation, cloud, and observability systems across all marketplace teams.
Build and own the evaluation pipeline with golden tasks, graders, regression tests, and release gates that make agentic workflow correctness measurable and cheap enough that teams actually use it.
Implement observability and cost control including traces, telemetry, token usage, cost-per-workflow, rate limits, and failure handling so stakeholders always know what is running, what it costs, and what to do when it breaks.
Create reusable scaffolding including templates, starter kits, wrapper scripts, and workflow automation that turn proven patterns from marketplace pilots into platform components every team can adopt in a day.
Partner with the AI Architect to make the reference architecture real, in code, in production, on a schedule that does not slip.
What you will bring:
A track record of shipping internal developer platforms or developer tooling that real teams depend on, with users, SLAs, and on-call rotations, not just side projects.
Hands-on experience designing and operating secure execution environments including sandboxing, runtime isolation, RBAC, secrets, and audit logging at production scale.
A history of shipping integrations across source control, CI/CD, ticketing, documentation, cloud, and observability systems across polyglot estates.
Experience building or operating tool gateways, adapters, or MCP-style integrations for agents and an understanding of the new failure modes they introduce.
An instrumentation-first mindset where traces, telemetry, cost-per-workflow, and eval pipelines are standard practice, not afterthoughts.
An SRE mindset where reliability, rollback, rate limiting, graceful degradation, and cost control are first-class concerns from day one.
A record of shipping reusable templates and starter kits that teams actually adopt rather than internal frameworks that gather dust.
Comfort across modern infrastructure including containers, cloud (AWS or equivalent), IaC, CI/CD, APIs, and scripting languages.
Experience supporting many engineers or multiple teams as internal customers, balancing opinionated defaults with inevitable special cases.
The ability to collaborate closely with security, legal, and risk stakeholders without losing momentum.
Nice to have:
Prior experience building agentic platforms or AI infrastructure at multi-team scale.
Familiarity with MCP (Model Context Protocol) integrations or equivalent agent tool gateway patterns.
Experience with multi-cloud or multi-stack platform engineering across distributed marketplace or e-commerce environments.
Exposure to cost governance tooling, token budgeting, or LLM observability platforms.
What we offer:
100% Remote Work
WFH allowance: Monthly payment as financial support for remote working.
Career Growth: We have established a career development program accessible for all employees with a 360º feedback that will help us to guide you in your career progression.
Training: For Tech training at Zartis, you have time allocated during the week at your disposal. You can request from a variety of options, such as online courses (from Pluralsight and Educative.io, for example), English classes, books, conferences, and events.
Mentoring Program: You can become a mentor in Zartis or you can receive mentorship, or both.
Zartis Wellbeing Hub (Kara Connect): A platform that provides sessions with a range of specialists, including mental health professionals, nutritionists, physiotherapists, fitness coaches, and webinars with such professionals as well.
Multicultural working environment: We organize tech events, webinars, parties, and activities to do online team-building games and contests.
Annual salary information is not provided for this position. Explore salary ranges for similar roles in our Salary Directory ›
This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.







